Papers with Switchboard corpus

5 papers
Expectation and Locality Effects in the Prediction of Disfluent Fillers and Repairs in English Speech (N19-3)

Copied to clipboard

Challenge: a study aims to understand the role of disfluencies in speech production . speakers tend to lessen cognitive load for upcoming difficulties .
Approach: They examine the role of three influential theories of language processing in predicting disfluencies in speech production.
Outcome: The proposed classifiers predict disfluencies in English conversational speech . the classifier features lexical surprisal, word duration and DLT integration costs .
Planning and Generating Natural and Diverse Disfluent Texts as Augmentation for Disfluency Detection (2020.emnlp-main)

Copied to clipboard

Challenge: Existing approaches to disfluency detection heavily depend on labeled data.
Approach: They propose a Planner-Generator based disfluency generation model that generates natural disfluent texts as augmented data.
Outcome: The proposed model outperforms baselines and leads to state-of-the-art performance on Switchboard corpus.
Annotation of Communicative Functions of Short Feedback Tokens in Switchboard (2022.lrec-1)

Copied to clipboard

Challenge: lexical forms and prosodic characteristics of short feedback tokens are not indicative of their communicative function.
Approach: They propose to annotate short feedback tokens with a lexical annotation scheme . they find that feedback functions have distinguishable prosodic characteristics .
Outcome: The proposed annotations show that lexical forms alone are not indicative of the communicative function.
Topic Spotting using Hierarchical Networks with Self Attention (N19-1)

Copied to clipboard

Challenge: Existing systems struggle to have consistent long term conversations with the users and fail to build rapport.
Approach: They propose a hierarchical model with self attention for topic spotting . they compare it to previous proposed techniques for topic detection .
Outcome: The proposed model outperforms existing models for topic spotting and deep models for text classification in an online setting.
Variation in Coreference Strategies across Genres and Production Media (2020.coling-main)

Copied to clipboard

Challenge: a lack of work on automatic coreference resolution on spoken and written language has led to inconclusive results.
Approach: They propose to use Ontonotes, Switchboard and Twitter to investigate coreference . they find fairly clear patterns of "behavior" for the different genres/medias .
Outcome: The results show that the choice of genre and the medium (spoken versus spoken) relates to the spokenwritten spectrum for coreference strategies.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations